Papers with text-based reward models

1 papers
Transferring Textual Preferences to Vision-Language Understanding through Model Merging (2025.acl-short)

Copied to clipboard

Challenge: Large vision-language models (LVLMs) perform outstandingly across multimodal tasks, but training them with preference data is computationally expensive.
Approach: They propose to merge text-based reward models with LVLMs to create visionlanguage reward models (VLRMs) this approach offers an efficient method for incorporating textual preferences into LVRMs.
Outcome: The proposed model improves over LVLMs’ scoring and text-based RMs, and offers an efficient method for incorporating textual preferences into LVRMs.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations